Brain and Language
○ Elsevier BV
Preprints posted in the last 30 days, ranked by how well they match Brain and Language's content profile, based on 11 papers previously published here. The average preprint has a 0.01% match score for this journal, so anything above that is already an above-average fit.
Hooper, J.; Dengler, J.; Basilico, D.; Nelson, M. J.
Show abstract
Sentence comprehension requires the incremental construction of syntactic structure and semantic interpretation. Prior neural work (Nelson et al., 2017) identified key neural events at major phrase boundaries during sentence comprehension. To investigate a behavioral correlation of these processes, we used self-paced reading to examine the impact of syntactic phase boundaries, semantic congruence, and sentence structure on sentence processing. Participants read object-relative, subject-relative, and canonical control sentences one word at a time and a subsequent comprehension task. Reading times were analyzed relative to phrase boundaries, node-closing operations, and semantic congruence. Object-relative sentences produced the greatest processing difficulty, demonstrated by increased reading times and decreased comprehension accuracy. Reading times peaked at the phrase boundaries, indicating that processing costs are tied to constituent completion rather than individual lexical categories. Reading times also increased with the number of syntactic constituents completed at a phrase boundary. Agent-patient semantic congruence produced its largest effects in object-relative sentences, suggesting that semantic information interacts with syntactic computations when processing demands are greatest. These findings demonstrate that self-paced reading is sensitive to the incremental processing associated with syntactic constituent completion. Processing costs are tied more closely to phrase completion than to individual lexical categories, scale with the amount of syntactic structure completed at a boundary and interact with agent-patient semantic interpretation during object-relative sentence comprehension. Together, these findings support a view of sentence comprehension in which syntactic structure building and semantic interpretation proceed incrementally and interact continuously throughout online language processing.
Andrade, K. D.; Melton, D. L.; Ries, S. K.
Show abstract
Language production requires the coordination of multiple cognitive processes. The ability to anticipate and override a habitual response in favor of a contextually-appropriate response are key subprocesses of cognitive control which enable speakers to communicate effectively. Word retrieval involves the co-activation of semantically related alternatives from which the speaker must select the appropriate target representation. Although cognitive control mechanisms have been proposed to contribute to resolving semantic interference during language production, the nature of these control processes remain unclear. Studies investigating the temporal dynamics of cognitive control during decision making tasks have led to a distinction between two operating processes: proactive control, initiated prior to the occurrence of conflict, and reactive control recruited after conflict is detected. We investigated the roles of proactive and reactive control in resolving interference between competing linguistic representations during word retrieval. We analyzed congruency sequence effects combined with delta-plot distributional analyses to dissociate potential adjustments in proactive versus reactive cognitive control in a picture-naming task manipulating semantic context compared to a minimally-linguistic Stroop-like paradigm. Reaction time distributional properties following semantically related trials revealed the engagement of proactive control in semantic interference resolution during word retrieval in the PWI task. In contrast, reactive inhibitory control was engaged in resolving semantic interference following low conflict trials. This distinction was not present in the minimally-linguistic task, which did not appear to engage adaptive control to the same extent. These findings demonstrate that both proactive and reactive cognitive control mechanisms contribute to language production, and are engaged dynamically, adjusting trial-by-trial to resolve semantic interference during word retrieval. In addition, our study provides important insight into the comparison of language with other cognitive domains and positions linguistic paradigms as being instrumental in the study of cognitive control dynamics.
XU, M.; REN, Y.
Show abstract
Building upon foundational psychological theories of event segmentation, this study addresses the limitation of overreliance on temporal boundaries as the primary segmentation criterion. Drawing on two experiments of direct and indirect causation in Mandarin Chinese, this study demonstrates how cognitive segmentation granularity and semantic integration jointly shape syntactic encoding. Results reveal distinct event encoding patterns for direct and indirect causation: coarse-grained segmentation leads to compact syntactic structures (e.g., verb-resultatives), while fine-grained segmentation yields varied multi-clausal expressions. Chinese speakers update event models via prediction errors of intentionality and protagonists, and tend to establish event boundaries at goal-relevant action endpoints when construing causal chains. These conceptual dimensions exert a modulating influence on both event segmentation and semantic integration. We propose a triad model integrating event segmentation, semantic integration, and linguistic specificity, providing a unified framework for elucidating the mind-language interface in conceptual construction and event coding of causation.
Xin, Y.; Xu, H.; Cong, F.; He, W.; zhang, g.
Show abstract
Audiovisual semantic matching can be achieved using either written words or pictures, yet whether these formats engage shared semantic matching representations with similar temporal dynamics remains unclear. We recorded electroencephalography from 27 participants while they performed audiovisual semantic matching tasks in which spoken words were paired with either written words or pictures. Stimuli included both natural and man-made objects. Time-resolved multivariate pattern analyses (MVPA or decoding), cross-decoding, and temporal generalization analyses were used to characterize the temporal dynamics of semantic processing. Reliable decoding of matching versus mismatching judgments emerged in both word and picture conditions. Decoding onset that significant above chance level occurred earlier for written words than for pictures and cross-decoding analyses revealed successful generalization between word and picture formats. Temporal generalization analyses further demonstrated distinct representational dynamics across formats, with word processing characterized by predominantly time-specific neural representations and picture processing showing more sustained and temporally stable representations. In addition, matching-related discrimination emerged earlier for natural objects than for man-made objects across both formats. The results suggest that speech-word matching shows earlier neural evidence of audiovisual alignment than speech-picture matching, potentially reflecting differences in how auditory linguistic input is integrated with visual information across representational formats.
Bahar, N.; Arabadzhiyska, D.; Jones, H.; Singh, S.; Davis, M.; Ricketts, J.; Ripolles, P.; Krishnan, S.
Show abstract
Contextual word learning is a fundamental mechanism for vocabulary acquisition during childhood. In adults, successful inference of word meaning from context is intrinsically rewarding, and is associated with greater enjoyment and greater activity in reward-related brain regions. Whether similar reward mechanisms support word learning in children, and whether they differ as a function of ability, remains unknown. We used functional magnetic resonance imaging (fMRI) to examine neural responses during contextual word learning in 25 children aged 11-13 years with typical reading skills and in 20 age-matched children with dyslexia. Neurotypical readers showed enhanced activation in core reward-processing regions, including the ventral striatum, when successfully learning the meanings of novel words. In contrast, children with dyslexia did not exhibit comparable reward-related responses despite performing the same task. Crucially, this group difference was specific to word learning, as no significant group differences were observed in ventral striatal responses during a non-linguistic monetary reward task. In addition, to confirm the behavioural relevance of these neural findings, we examined an age-matched, independent sample of children. We found that stronger reading skills were associated with greater enjoyment during successful word learning. Together, these results suggest that interactions between reward and language systems during contextual word learning is influenced by reading proficiency. Reduced intrinsic reward responses to successful language learning may contribute to differences in reading development and have implications for the design of more engaging and effective reading interventions for struggling readers.
Rizzi, R.; Stirn, J. R.; Eisenhut, Z.; Bidelman, G. M.
Show abstract
Listeners discretize the speech signal by assigning sounds to phonetic categories, though there is variability in how individuals accomplish categorization. Having more consistent categorization of sounds may be advantageous for understanding speech-in-noise (SIN). Though, it is unclear how different levels of neural processing in the auditory system reflect these perceptual differences. We recorded brainstem frequency-following responses (FFRs) and cortical event-related potentials (ERPs) while listeners actively labeled vowels along an acoustic-phonetic continuum using a visual analog scale. We computed intertrial consistency of neural responses to index the stability of listeners' neural speech representations across stimulus presentations. We also assessed how faithfully midbrain and cortical responses represented stimulus acoustics using representational dissimilarity matrices (RDMs) computed across all token pairs. Neural RDMs were then compared with acoustic and phonetic category RDMs to assess whether FFRs and ERPs carried gradient vs. categorical information of the speech signal. We found greater behavioral consistency during phoneme labeling was correlated with improved SIN scores. Neurally, we found greater cortical or subcortical consistency predicted greater behavioral consistency. RDMs revealed subcortical responses retained more acoustic details, while cortical responses more closely reflected abstract phoneme categories. Our findings reveal important benefits of perceptual consistency to other domains of speech perception. We find perceptual consistency is driven by more consistent encoding of speech at either a cortical or subcortical level. More consistent sensory processing could provide a more stable readout of the speech signal to higher cortical brain areas which could confer advantages to later perceptual processes downstream.
Rizzi, R.; Stirn, J. R.; Eisenhut, Z.; Bidelman, G. M.
Show abstract
Successful speech perception requires listeners to bin continuous acoustic information into discrete phonetic categories. However, some people maintain within-category acoustic information (gradient) while others discard category-irrelevant information (discrete) during perception. Listeners also vary in how consistently they label speech sounds and more gradient/consistent labeling has been linked with better speech-in-noise (SIN) perception. Here, we test how neuroanatomical properties of the brain's major speech-language and auditory pathways relate to individual differences in speech categorization and SIN processing. We measured phonetic categorization and SIN comprehension via phoneme labeling and QuickSIN tasks. Diffusion-weighted imaging (DWI) with probabilistic tractography estimated axonal density within the bilateral arcuate fasciculi and brainstem-cortical auditory projections. Anatomical morphology (surface area, gray matter volume, thickness) was also quantified in the adjacent frontotemporal cortical areas and midbrain. Behaviorally, we found more consistent categorizers had better performance on the QuickSIN. DWI showed that more gradient listeners had greater white matter density in the left arcuate fasciculus and brainstem-cortical auditory pathways, while better SIN performance was predicted by denser white matter in the brainstem-cortical auditory pathways. Morphometric results revealed more consistent listening was associated with greater cortical thickness in right superior temporal gyrus and more gradient listening was associated with greater surface area in right pars opercularis. We infer that individual differences in phonetic categorization relate to SIN comprehension and are at least partially explained by neuroanatomical properties of the auditory-linguistic brain.
BAEK, S.-C.; Kim, S.-G.; Maess, B.; Grigutsch, M.; Sammler, D.
Show abstract
Prosody is a fundamental aspect of speech characterized by suprasegmental features such as pitch. Prosodic pitch contours are used to convey speakers intentions, for example, to make a statement or ask a question. Understanding these intentions requires abstracting continuous, variable pitch information into discrete categories. Category perception has been proposed to recruit the motor system in an effector-specific manner, whereby cortical areas controlling motor effectors support speech sound recognition by identifying articulatory gestures. However, it remains unclear whether effectors involved in pitch production similarly contribute to prosodic category perception. To address this question, we collected magnetoencephalography data from 29 participants (15 females) while they first sang pitches arranged in five-tone melodies and then identified the prosody (Statement vs. Question) of single words varying in pitch contour along a five-level continuum. Using a region-restricted searchlight approach to decode singing from rest, we localized two premotor regions for pitch production, corresponding to the ventral and dorsal laryngeal motor cortex (LMC). A separate neural decoding analysis revealed that perceived prosodic categories were decodable in these regions, especially from the dorsal LMC that is more closely associated with pitch regulation. Importantly, decoding performance mirrored behavioral discriminability of prosodic categories across the continuum, suggesting that these regions are involved in perceptual decision-making. Finally, pitch motor areas exchanged category-related information with auditory regions, indicating these areas do not merely echo the processing in auditory regions. Together, these findings highlight effector-specific motor support for prosodic category perception, thereby broadening our understanding of motor involvement in speech perception. Significance StatementSpeech perception has been proposed to recruit the premotor cortex, with different subregions linking speech sounds to the articulatory gestures used to produce them. We investigated this idea through prosody--pitch changes in speech conveying meanings such as statements and questions. Using magnetoencephalography, we identified pitch motor areas during a singing task and tested whether they represent perceived prosodic categories. We found that prosodic categories were distinguishable in these regions and that this neural discriminability mirrored behavioral discriminability across clear and ambiguous prosody, suggesting involvement in perceptual decision-making. These findings are unlikely to reflect passive echoes from auditory regions, as pitch motor areas actively influenced them during categorical processing. Our results highlight effector-specific motor support for forming abstract prosodic representations.
Herrmann, B.; Fink, L. K.; Pandey, P. R.; Johnsrude, I.; Ryan, J. D.
Show abstract
Speech comprehension in noisy environments often requires cognitive effort, but listeners may disengage when comprehension becomes impossible. Eye movements have recently emerged as a promising new measure of listening effort, but it remains unclear whether eye movements are sensitive to the full effort profile across easy, difficult, and impossible speech comprehension. Across four experiments, participants listened to sentences at easy, difficult, and impossible levels of multi-talker background babble while pupil size and eye movements were recorded. Pupil size generally followed the expected inverted u-shaped effort profile: low for easy speech, maximal for difficult but still intelligible speech and lower again for impossible speech, although this pattern partly reflected sustained, condition-specific differences and not only sentence-evoked responses. Gaze dispersion - measuring the spread of eye movements - decreased with high temporal selectivity during difficult relative to easy and impossible speech, indicating reduced eye movements during active, effortful listening. However, gaze dispersion was also lower, but less temporally selective, during impossible compared to easy listening, especially in non-baseline-corrected analyses, suggesting that reduced eye movements do not index listening effort uniquely. Instead, eye movements appear to reflect both attentional engagement during difficult listening and disengagement or inward attention when meaningful listening is no longer possible. These findings indicate that pupil size and eye movements provide complementary indices of listening-related cognition, and highlight the integration of listening, cognition, and motor systems.
Ji, Y.; Qian, Y.; Wang, Y.; Li, J.; Li, Y.; Lin, W.; Bi, H.-Y.; Zhang, P.
Show abstract
While evidence suggests magnocellular deficits in the geniculostriate pathway in adults with dyslexia, neural deficits in the subcortical pathways during childhood remain unclear. Here, we used high-resolution fMRI to investigate subcortical abnormalities in Chinese children with developmental dyslexia. Fast achromatic motion stimuli and slowly drifting chromatic gratings were used to assess magnocellular (M) and parvocellular (P) functions, respectively. Relative to controls, children with dyslexia showed a selective reduction in responses to the M stimulus in the ventromedial pulvinar (vmPul) and the superficial layers of the superior colliculus (SCs), along with significantly reduced SCs-vmPul connectivity. Importantly, while vmPul responses to the M stimulus were positively associated with reading skills in healthy controls, this correlation was absent in children with dyslexia. Unlike previous findings in adults, the lateral geniculate nucleus (LGN) exhibited a non-selective reduction in responses to both stimuli, no volume reduction, and no correlation with reading ability. These findings demonstrate a selective deficit to achromatic motion processing in the colliculus-pulvinar pathway in children with dyslexia, which contributes to their reading difficulties. This early subcortical disruption differs from, and precedes, the neural deficits previously reported in the adult LGN, offering new insight into the developmental trajectory of dyslexia.
Pesthy, O.; Toth-Faber, E.; Nagy, C.; Nemeth, M.; Janacsek, K.; Nemeth, D.
Show abstract
Children often outperform adults in probabilistic statistical learning tasks, yet the mechanisms underlying this developmental advantage remain poorly understood. Here, we used eye-tracking measures of belief updating to examine how children and adults acquire and update predictions in a probabilistic sequence-learning task. Using the standard (oculomotor) reaction time measure, children showed stronger statistical learning than adults, replicating previous behavioral findings while revealing a more detailed profile of developmental differences in statistical learning. Critically, children updated their predictions more frequently: they were less likely to repeat previous predictions and more likely to shift their expectations in response to new input. Adults, in contrast, showed greater persistence, tending to maintain prior predictions even when those predictions were inconsistent with the underlying statistical structure. Despite these pronounced differences in updating behavior, the processing and use of prediction errors were remarkably similar across age groups. These findings indicate that developmental differences in statistical learning do not primarily arise from how prediction errors are computed, but rather from how prior beliefs and incoming information are weighted during belief updating. Children's enhanced learning may therefore reflect reduced reliance on stable priors and greater sensitivity to current sensory evidence, supporting a more exploratory learning strategy. Adults, by contrast, appear to favor an exploitative strategy that stabilizes existing predictions but reduces flexibility in probabilistic environments. More broadly, the results suggest that developmental changes in statistical learning may reflect age-related differences in how readily learners revise their predictions in response to incoming evidence. By integrating sensitive oculomotor measures with analyses that probe the mechanisms underlying belief updating, the present study provides a more fine-grained account of how predictive learning changes across development and offers a framework for reconciling previously inconsistent developmental findings in statistical learning.
Galeano-Otalvaro, J.-D.; Dieudonne, B.; Francart, T.; Wouters, J.
Show abstract
Understanding speech in noisy environments relies strongly on binaural cues such as interaural time differences (ITDs) and interaural level differences (ILDs), which support spatial hearing and the segregation of competing sound sources. When these cues are degraded, listeners experience substantial difficulty in complex acoustic environments. Behavioural measures of binaural benefit, such as binaural masking level differences (BMLDs), binaural intelligibility level differences (BILDs), and spatial release from masking (SRM), are well established in normal-hearing (NH) listeners, but they require an active behavioural response. Neural speech tracking using electroencephalography (EEG) has emerged as a promising approach for quantifying neural processing of continuous speech, yet its sensitivity to spatial hearing cues remains insufficiently characterised. In this study, we investigated the neural correlates of spatial release from masking in NH listeners using EEG-based neural speech tracking. Nineteen participants listened to continuous Dutch speech stories presented with masking noise under two spatial configurations, collocated (S0N0) and spatially separated (S0N90), across multiple signal-to-noise ratios (SNRs). Neural tracking of the speech envelope was quantified using both envelope reconstruction and temporal response function (TRF) analyses. Spatial separation enhanced neural tracking of the target speech envelope, particularly at challenging SNRs where behavioural SRM was also observed. TRF analysis further revealed condition-dependent morphologies, including increased amplitudes and decreased latencies of late cortical components consistent with spatial unmasking effects. These neural differences were most pronounced at low SNRs, where spatial cues provide the greatest perceptual benefit. Together, these findings demonstrate that neural speech tracking captures cortical signatures of spatial unmasking and closely reflects behavioural improvements in speech understanding. Establishing these relationships in NH listeners supports the development of objective neural measures for evaluating binaural benefit in difficult-to-test populations.
Lau, J. C. Y.; McHaney, J. R.; Goldman, L.; Robinshaw, K.; Mou, F.; McFarlane, K.; Chandrasekaran, B.; Losh, M.
Show abstract
Reported perceptual differences in autism may arise from reduced use of prior context to shape incoming sensory input. Speech perception provides a critical test of this account because stable perception requires listeners to integrate variable acoustic signals with contextual expectations. This study examined context-dependent modulation of speech encoding in autistic and non-autistic adults using the frequency-following response (FFR), a neurophysiological measure of phase-locked auditory encoding. Participants heard English intonational pitch contours presented in repetitive and variable contexts while EEG was recorded. Principal component analysis of FFR metrics yielded components indexing neural encoding fidelity and timing. Non-autistic participants showed enhanced encoding fidelity in more predictable contexts, whereas autistic participants showed reduced context-dependent modulation. Neural encoding timing also showed divergent context effects across groups, suggesting altered balance between feedback-based predictive mechanisms and locally driven adaptation processes. Within the autistic group, greater context-related modulation of encoding fidelity was associated with lower ADOS-2 Social Affect severity but poorer speech-in-noise perception, suggesting that the functional impact of contextual modulation depends on input reliability and task demands. These findings indicate that context-dependent modulation of speech encoding is altered in autism and may contribute to individual differences in auditory and social-communicative function.
Li, J.; Hiersche, K.; Aryeetey, N.-A.; Quatrale, A.; Resnick, P.; Saygin, Z. M.
Show abstract
The visual word form area (VWFA) is a hallmark of literacy in the human brain. Alongside reading acquisition, the functional organization of the ventral temporal cortex (VTC) undergoes substantial changes. However, it remains unclear what factors uniquely drive the emergence and continued development of the VWFA, or more broadly shaping category selectivity across the VTC. Here, we combined cross-sectional and longitudinal data from children in early childhood (3-9 years) to investigate the development of visual language selectivity. We found that after controlling for age, literacy acquisition drove increases in word selectivity of the VWFA, but not the continued development of other early-developed category selectivity. Sensitivity to spoken higher-level language information within the VWFA was also associated with reading ability. By projecting the VWFA defined at the later time point onto each childs earlier time point, we also found that the pre-VWFA showed no preferential tuning to any particular visual category. Finally, longitudinal changes within the VWFA, both increased responses to visually presented words and decreased responses to auditory control conditions, were associated with changes in functional connectivity of VWFA to high-level language regions, even after controlling for initial activity levels. Together, the current study shows a unique role for literacy acquisition and experience-dependent connectivity changes in the emergence and functional specialization of the VWFA, providing empirical evidence for the revised neuronal recycling hypothesis and connectivity hypothesis of functional brain organization.
Husta, C.; Seijdel, N.; Drijvers, L.
Show abstract
Face-to-face communication requires listeners to attend, integrate, and weigh multiple communicative signals, including auditory speech, mouth movements, and co-speech gestures. The contribution of these signals may depend on the reliability of auditory input and the informativeness of the available signals. We utilized rapid invisible frequency tagging (RIFT) with EEG to examine how participants attend to and integrate these different signals in clear and adverse listening conditions. Participants watched videos of an actress producing clear or noise-vocoded sentences. Auditory speech was amplitude-modulated at 58Hz, while the luminance of the gesture and mouth regions was frequency-tagged at 63Hz and 65Hz. Degraded speech elicited stronger responses at the auditory tagged frequency, suggesting increased attentional gain to the auditory signal when listening was challenging. In contrast, clear speech elicited stronger responses at the gesture tagged frequency and a stronger 2Hz intermodulation response (65-63Hz), reflecting enhanced nonlinear coupling between mouth movements and gestures. Finally, in degraded speech, the informativeness of mouth movement, but not gesture, was associated with intermodulation strength, suggesting that the informativeness of mouth movements plays a greater role in multisensory interaction when listening is challenging. Our findings demonstrate that both signal reliability and informativeness shape multisensory integration during spoken language comprehension.
Nidiffer, A.; O'Sullivan, A.; Lalor, E. C.
Show abstract
In noisy environments, visible speech articulations improve listening comprehension. The benefit derives from several sources, including articulatory timing and shape. Recent research has shown that visual cortex encodes a categorical representation of articulatory features and that visual speech can benefit both acoustic and phonetic feature processing separately. The present study advances the hypothesis that the shape of the articulators specifically influences the categorization of auditory speech in terms of its phonetic features. We tested this by linearly modeling electroencephalographic responses to natural, continuous speech (in noise) in terms of the acoustic and articulatory features of the speech. We compared the performance of these models in conditions where the speech was accompanied by a natural video of the speaker with their mouth visible, and a video where their mouth was covered by a dynamic ellipse obscuring articulatory shape but preserving dynamics. The dynamic mask reduced comprehension, neural processing of phonetic features, the associated multisensory benefits, and indices of visual-only linguistic processing over occipital scalp. Our findings support substantial visual involvement in speech comprehension, derived largely from the shape of the articulators. They also corroborate several proposals involving audiovisual speech processing hierarchy and the nature of the information contained in visible speech. HighlightsO_LIVisual speech provides at least two forms of information to enhance acoustic speech processing: redundant temporal dynamics and complementary articulatory information C_LIO_LICovering the mouth with a dynamic mask preserves horizontal and vertical lip movement information, but largely removes articulatory detail C_LIO_LIVisual speech with a mask preserves some general multisensory benefits but removes visual linguistic information and its ability to enhance auditory processing at the level of phonetic features. C_LI
Chiyohara, S.; Asai, T.; Hiromitsu, K.; Imamizu, H.
Show abstract
Working memory (WM) is a core cognitive function that supports goal-directed behavior by temporarily maintaining and manipulating information. One of the most widely used paradigms for investigating WM function is the N-back task, and numerous neuroimaging studies have examined load-dependent neural responses using a variety of analytical approaches. However, most previous studies have focused on low-to-moderate load ranges (primarily 0-3-back), and it remains unclear how whole-brain activity patterns reconfigure across a broader range of WM demands, including conditions approaching capacity limits. In the present study, we investigated behavioral performance and whole-brain activity patterns across an extended N-back task ranging from 0-back to 7-back. Behavioral analyses revealed that discrimination sensitivity (d') decreased nonlinearly with increasing WM load, whereas reaction time (RT) exhibited an inverted-U pattern, peaking at intermediate load conditions. To characterize load-dependent whole-brain activity patterns, we computed relative activation maps by subtracting the participant-wise mean activation map across all conditions from each condition-specific activation map. Spatial similarity analyses with the Yeo 7-network templates revealed that low-load conditions showed relatively high similarity to default mode network (DMN)-related patterns. Similarity to the dorsal attention network (DAN) and frontoparietal network (FPN) was maximal at intermediate load levels, indicating load-dependent changes in network similarity profiles. High-load conditions were characterized by partial re-emergence of DMN-related patterns, accompanied by reduced DAN/FPN similarity. In addition, semantic similarity analysis using Neurosynth-derived semantic maps revealed relatively high similarity to default mode-related and self-referential representations under low-load conditions. Intermediate-load conditions showed strong correspondence with working memory- and executive control-related representations, whereas high-load conditions exhibited increased similarity to salience-, aversive/interoceptive-, and inhibitory-control-related representations. Together, these findings suggest that increasing WM load is associated not merely with stronger activation, but with changes in whole-brain activity patterns accompanied by nonlinear changes in network similarity profiles across levels of cognitive demand. Furthermore, the relative activation map-based whole-brain pattern analysis used in this study may provide a useful approach for evaluating changes in whole-brain state representations associated with cognitive load.
Zhang, X.; Li, Z.; Zhang, D.
Show abstract
Understanding speech in noise is a central challenge of everyday communication, yet listeners often succeed by using prior context. How the brain uses such context remains debated: it may refine predictions about upcoming words, or it may provide a higher-level framework that helps degraded speech cohere into meaning. Here we combined simultaneous EEG-fNIRS recording with hierarchical multivariate encoding models to track how prior context shapes speech processing from acoustics to words and sentential meaning. Participants listened to natural spoken narratives under clear speech, noisy speech, and context-supported noisy speech conditions. Context brought comprehension of noisy speech close to clear-speech levels. EEG revealed that contextual support reduced neural encoding of lexical surprisal and entropy, indicating weaker tracking of local word-level prediction demands. In contrast, when context was available, fNIRS showed enhanced encoding of sentence-level semantic integration across frontal regions and the right angular gyrus, and stronger angular gyrus encoding predicted better comprehension. By combining EEG and fNIRS to capture complementary electrophysiological and hemodynamic signals, this multimodal approach reveals a hierarchical shift in degraded speech comprehension: prior context does not simply improve word-by-word prediction, but scaffolds the integration of noisy input into coherent discourse.
Klis, A.;Menn, K.;Cetincelik, M.;Snijders, T.;Junge, C.
Show abstract
Speech consists of regularities at different timescales. Already during infancy, neural electrophysiological activity aligns to these rhythms. The degree to which infants exhibit neural tracking of speech can be linked to their language development. In this study, we examined how the neural tracking of sung speech develops across age, from infancy to early childhood, and across different frequency bands (i.e., at the stress, syllabic, and phonemic rates), and whether neural tracking at each frequency and age predicts childrens language outcomes. We included 2565 children of the longitudinal YOUth cohort. Children listened to Dutch sung nursery rhymes while EEG was recorded at three measurement waves. After preprocessing the data, we included 955 children at 5 months, 1048 children at 10 months, and 795 children at 2-4 years. The final sample consisted of 750 children who also completed a receptive vocabulary test at 2-4 years. Children from 5 months onwards showed significant neural tracking of stressed syllables, syllables, and phonemes, measured with speech-brain coherence (SBC). Unexpectedly, there were no developmental changes in SBC across different frequency bands from infancy to early childhood. As expected, children with larger receptive vocabularies showed increased SBC in the stressed syllable rate. These findings suggest that stronger tracking of stressed syllables is related to individual differences in language ability.
Messi, A.-P.; Bhuyain, A.; Pylkkänen, L.
Show abstract
How the brain constructs meaning across extended contexts remains poorly understood. While neural responses to words and sentences are well characterized, much less is known about the brain mechanisms supporting narrative comprehension. Sentence-level studies suggest that neural activation increases as word meanings are integrated into sentence meaning. At the discourse level, theories propose that narratives depend on situation models, possibly engaging networks beyond core language regions, including the default mode network. Because narrative comprehension unfolds over longer timescales, processing time may be a bottleneck. In this MEG study, we tested how representation size and presentation rate shape neural responses by varying linguistic structure (words, sentences, stories) and the speed of visual text in 1-4-word chunks. We found an early bilateral story effect in visual cortex, followed by a spatiotemporal progression of activity along the temporal lobes that culminated in a three-way contrast among word lists, sentence lists, and stories. Faster presentation altered this pattern: the left-lateralized story effect disappeared, and the right-lateralized effect became more spatially restricted. Under Fast presentation, significant effects were limited to left lateral language cortex distinguishing coherent inputs from word lists, and to two right-hemisphere story effects in extended language regions. We also observed a context effect in the Slow Story condition, with neural responses remaining constant as the narrative unfolded while they increased in the SentenceList and WordList conditions. This effect was absent under Fast presentation, suggesting story-specific comprehension that is temporally constrained. Together, the findings identify temporal constraints as a key determinant of the neural signatures of narrative processing.